Data Engineering Startups funded by Y Combinator (YC) 2026

September 2026

Browse 88 of the top Data Engineering startups funded by Y Combinator.

We also have a Startup Directory where you can search through over 5,000 companies.

  • Fivetran
    Fivetran
    Y Combinator LogoW2013
    Active • 1,200 employees • San Francisco
    Fivetran automates data movement out of, into and across cloud data platforms. We automate the most time-consuming parts of the ELT process from extracts to schema drift handling to transformations, so data engineers can focus on higher-impact projects with total pipeline peace of mind. With 99.9% uptime and self-healing pipelines, Fivetran enables hundreds of leading brands across the globe, including Autodesk, Conagra Brands, JetBlue, Lionsgate, Morgan Stanley, and Ziff Davis, to accelerate data-driven decisions and drive business growth. Fivetran is headquartered in Oakland, California, with offices around the world. 
    data-engineering
    saas
    analytics
    b2b
  • Magma
    Magma
    Y Combinator LogoS2026
    Active • 1 employees • San Francisco
    We enable companies monetize their agents' traces.
    ai
    reinforcement-learning
    data-engineering
  • Hebbian Robotics
    Hebbian Robotics
    Y Combinator LogoS2026
    Active • 2 employees • San Francisco
    Hebbian Robotics builds APIs for robotics data teams to verify data quality for model training. Our customers use us to build quality control pipelines in just a few lines of code, with custom models 10x faster and 8x cheaper than general foundation models. https://hebbianrobotics.com/ https://github.com/Hebbian-Robotics/hflow
    robotics
    data-engineering
    data-science
    databases
    ai
  • Olam Labs
    Olam Labs
    Y Combinator LogoS2026
    Active • 2 employees • San Francisco
    Multi-agent environments let us evaluate and train models in complex, simulated worlds. We work with researchers on both evaluations and training for character and agentic performance typically difficult to assess for using current popular datasets. Play one of our first releases, Multi-Agent Arena (https://olamlabs.ai/arena), where humans come play social strategy games against multiple AI agents. It's currently top 50 on OpenRouter and has thousands of matches played each day. We use Multi-Agent Arena to build datasets on agentic performance and real-world socialization.
    reinforcement-learning
    gaming
    data-engineering
    artificial-intelligence
  • Alkera AI
    Alkera AI
    Y Combinator LogoS2026
    Active • 3 employees • San Francisco
    Alkera is a data engineering/analysis/science agent that works in your IDE or CLI. If you’re tired of agents that produce false, or even dangerous, reports about your data, then Alkera is for you. Our agent is equipped with your data’s history and lineage, your team’s knowledge, and tailored plugins for data apps such as Snowflake and dbt. Simply install and connect your services to begin doing trusted agentic data work. We also offer VPC & on-prem deployment options for when data can't leave your boundary.
    data-science
    data-engineering
    developer-tools
    enterprise-software
    artificial-intelligence
  • Litmus
    Litmus
    Y Combinator LogoS2026
    Active • 3 employees • New York City
    Litmus is building the most accurate framework for evaluating and benchmarking human capability, starting with software. AI will compound small differences in human capability into increasingly large differences in what people can accomplish, while making existing static benchmarks obsolete. Software is already there: AI can hill-climb any output-based evaluation, while the ability to direct it is becoming the defining advantage. Litmus applies the same approach we already use for model evals to humans – creating world-like environments, and inspecting trajectory instead of just output. Every knowledge industry will soon face the same problem. We build Litmus to tell you what humans are capable of. We're already helping build frontier technical teams at Mercor, Composio, Neo Scholars, and more.
    recruiting
    data-engineering
    ai
  • Petrarch
    Petrarch
    Y Combinator LogoS2026
    Active • 3 employees • San Francisco
    Petrarch brings internal company data to frontier labs, starting with manufacturing and industrial companies. We source data including codebases, project files, and payments from bankruptcy courts, then de-identify and prepare the data for model training.
    b2b
    marketplace
    data-engineering
  • Markov
    Markov
    Y Combinator LogoS2026
    Active • 2 employees • San Francisco
    Markov builds expert computer-use datasets for frontier AI labs. Our datasets include screen recordings + mouse/keyboard actions of professionals using CAD, design and enterprise software. We work with the top AI labs and we’ve gotten more than 350k downloads on HuggingFace. Our long-term goal is to be the full stack platform for computer-use AI, starting with data and then the infrastructure to train and deploy computer-use models. Dev and Harish are technical co-founders. Dev studied aerospace engineering at IIT Madras and was the youngest team member at Sarvam - India’s top AI lab. Harish is an Emergent Ventures grantee and his prior robotics work was featured in top national publications.
    reinforcement-learning
    data-labeling
    data-engineering
  • Ashr
    Ashr
    Y Combinator LogoW2026
    Active • 2 employees • San Francisco
    Ashr Manifold is a self-contained model training, hosting, observability, and continual learning platform for purpose-built open-weight models. Startups can use Manifold to post-train and manage fleets of continuously improving open-weight models without having a dedicated research team or infrastructure. Ultimately, this reduces operating costs and allows a greater degree of control over proprietary company knowledge. Enterprises can use Manifold to encode institutional judgements and priorities into open-weight models, reduce token costs, and manage every level of the AI life cycle, from data curation, to model serving on weights and infrastructure they own.
    artificial-intelligence
    reinforcement-learning
    data-engineering
    generative-ai
    deep-learning
  • Corvera
    Corvera
    Y Combinator LogoW2026
    Active • 4 employees • San Francisco
    Corvera (YC W26) is the context layer for AI-native CPG brands. We help brands unlock the full potential of AI by making their unified data legible to any AI tool via MCP. As of mid-March 2026, we scaled from $0 to $33k in MRR in 4 weeks, are serving 12 brands, and are growing 130% week-on-week. With Corvera in place, brands can enable anyone in their organization to: - Rapidly deploy AI agents via tools such as Claude and ChatGPT, - Build dashboards and apps using Cursor and Lovable, and - Automate workflows across the entire business, from category management to supply chain. All in a secure, audit-logged, and access-controlled environment, without the need for a team of data engineers or AI specialists, in a simple plug-and-play solution. We're founded by Chris (2x Founder; ex-CEO at Better Nature; Forbes 30U30), Dirk (ex-Google Data & AI Lead), and Matthew (ex-Head of Product at Rosemark; Princeton MEng in CS). To-date, we’ve raised $6.2M and are backed by YC, firstminute capital, 6 Degrees Capital, 20VC, Rebel Fund, Duke Capital Partners, and more. Our ICP are omnichannel retail-focused CPG brands at the $100M+ revenue inflection point. Our mission is bold: to enable the next generation of unicorn CPG brands to be run by teams of less than 10 people. Our vision is even bolder: to be the system of intelligence for the $5.4T global CPG industry. Visit www.corvera.ai to learn more!
    b2b
    saas
    data-engineering
    consumer
    ai
  • Velum Labs
    Velum Labs
    Y Combinator LogoW2026
    Active • 2 employees • San Francisco
    Velum is the operating system for data quality. Velum automatically monitors and enforces data quality across a company's data stack, so bad data never reaches dashboards. We turn data quality from a manual task into infrastructure that runs itself. Data trust you can prove. From the pipeline to the boardroom.
    machine-learning
    data-engineering
  • Captain
    Captain
    Y Combinator LogoW2026
    Active • 2 employees • San Francisco
    Captain syncs complex, multimodal files and cloud storage from sources like S3, SharePoint, and Google Drive, and makes that knowledge searchable for your agents. Instead of stitching together parsers, embeddings, vector databases, rerankers, and retrieval infrastructure, Captain self-tunes the retrieval pipeline to your data and workloads. On benchmarks, this improves accuracy from roughly 78% with standard RAG to over 97% by tuning retrieval to the unique structure and content of your data. Captain is purpose-built for complex, regulated, and high-stakes file search workloads. Learn more at https://captain.dev
    data-engineering
    infrastructure
    b2b
    api
  • Velvet
    Velvet
    Y Combinator LogoF2025
    Active • 3 employees • San Francisco
    World model datasets and research
    artificial-intelligence
    data-engineering
    conversational-ai
    generative-ai
  • Hyperspell
    Hyperspell
    Y Combinator LogoF2025
    Active • 8 employees • San Francisco
    Hyperspell is your Company Brain. AI agents are brilliant and clueless. They ace any test and still have no idea how your company works. Hyperspell connects your tools and synthesizes documents and conversations into a live, permissioned context graph. Any agent can read from it and write back to it like a filesystem, with every fact traced to source.
    ai
    machine-learning
    saas
    data-engineering
  • DeepAware AI (Robotics Center of Silicon Valley)
    DeepAware AI (Robotics Center of Silicon Valley)
    Y Combinator LogoS2025
    Active • 4 employees • San Francisco
    DeepAware (Robotics Center of Silicon Valley, https://roboticscenter.ai/) is the fastest way for enterprises and researchers to get robots and robotics parts in the US — 72-hour delivery or Bay Area pickup. Beyond hardware, we help teams collect teleoperation data, build reinforcement learning environments, and deploy robots into production. Customers include AI labs, industrial operations, research teams, and event producers.
    supply-chain
    data-engineering
    robotics
    machine-learning
    artificial-intelligence
  • sieve
    sieve
    Y Combinator LogoP2025
    Active • 2 employees • New York City
    sieve solves data cleaning for hedge funds and investment firms by letting them get clean data in four lines of code. Currently, their data pipelines have conditions that raise for human review, which literally send an email to engineers with data that needs to be reviewed. We provide an API that integrates directly into their existing pipeline - instead of raising for human review, they can send all the same information to our API and get clean, high-quality data back. By using our AI agents built specifically for financial data collection, along with expert-in-the-loop review, we provide our clients with clean, validated data at a scale and level of quality that wasn't achievable before.
    apis
    investing
    data-engineering
  • Vision Lab
    Vision Lab
    Y Combinator LogoP2025
    Active • 11 employees • San Francisco
    We capture and structure real factory workflows at scale by combining first-person industrial video with SOP-level process knowledge. This enables robotics and AI labs to train on real production data, not just controlled lab environments.
    manufacturing
    ai
    robotics
    data-engineering
  • Nitrode
    Nitrode
    Y Combinator LogoW2025
    Active • 8 employees • San Francisco
    Nitrode is a research company focused on improving game development with AI. To advance game development, we believe AI models firstly need to be judged against robust evaluation frameworks that reflect the actual complexity of the field.
    data-engineering
    machine-learning
    b2b
    artificial-intelligence
  • Kilvin
    Kilvin
    Y Combinator LogoF2024
    Active • 3 employees • New York City
    Kilvin is an AI-native agency that builds custom software solutions for non-technical teams. With AI, the future of software is democratized. Every team deserves tools built exactly for them, not off-the-shelf SaaS products they need to adapt to. Kilvin makes it a reality, delivering a private, secure workspace of custom applications, built around your workflows and ready in days, not months. No engineers needed on your side. Just software that fits exactly how you work.
    developer-tools
    automation
    data-engineering
    productivity
    ai
  • Melder
    Melder
    Y Combinator LogoF2024
    Active • 2 employees • New York City
    Melder is an Excel add-in that brings AI functions and document support into your spreadsheets. Upload files directly into cells, use smart formulas like =GEN, and build automations—all without leaving Excel. Core features: - File-to-Sheet: Drop PDFs directly into cells, then reference them in formulas. - AI-Powered Functions: Write formulas like =GEN() or =EXTRACT() to summarize, classify, and analyze content. - Chat Assistant: Use our AI assistant to help build your sheets or answer questions from your data, live in the workbook. Business users use Melder to: - Accelerate diligence by extracting insights from data rooms - Review contracts by identifying key terms and clauses instantly - Run market research by pulling information from competitor websites - Synthesize transcripts by generating summaries from interviews and calls Melder brings the power of structured spreadsheet logic to the messy, unstructured data world—no coding needed.
    data-engineering
    generative-ai
    artificial-intelligence
  • Sensei
    Sensei
    Y Combinator LogoS2024
    Active • 2 employees • Boston
    Sensei helps robotics companies scale and outsource their training data collection. Our hardware platform enables the collection of human-demonstration data at a tenth of the cost and twice the speed of current teleop approaches. Our software platform acts like Scale AI for robotics data: a large network of paid human operators use our low-cost collection platform to fulfill data-generation requests.
    robotics
    hard-tech
    artificial-intelligence
    marketplace
    data-engineering
  • Overstand Labs
    Overstand Labs
    Y Combinator LogoW2025
    Active • 4 employees • New York City
    Overstand is a data lab that allows our customers to navigate any set of data in just a few minutes. *Enterprise*: For enterprises, we we unify Slack, email, calls, and operational data, then surface the signals that matter — customer needs, risks, and revenue opportunities hidden in everyday conversations. *Legal Firms*: For legal firms, we help them really quickly understand their entire discovery corpus (either before, or after document review), and quickly build out an initial case assessment and facts. Instead of waiting on reports or manual analysis, teams get immediate, evidence-backed answers from the data they already have. Overstand delivers clarity and leverage as your business scales.
    data-science
    data-engineering
    conversational-ai
    artificial-intelligence
    legaltech
  • Sharpe
    Sharpe
    Y Combinator LogoS2024
    Active • 3 employees • San Francisco
    Sharpe helps traders go from idea to profit in minutes with AI, bundling petabytes of market data with high-performance infrastructure.
    ai
    finance
    analytics
    data-engineering
    big-data
  • Zeit AI
    Zeit AI
    Y Combinator LogoS2024
    Active • 12 employees • Munich, Germany
    Ask a question in plain English and ZeitMind connects SAP, your ERP, CRM, HR systems and 600+ other sources, prepares the messy data, and turns it into the live applications your teams actually make decisions with. It is built for the finance, controlling, procurement, supply chain and sales teams at companies that run enterprise systems without an enterprise data team.
    artificial-intelligence
    analytics
    b2b
    saas
    data-engineering
  • Mica AI
    Mica AI
    Y Combinator LogoS2024
    Active • 3 employees • San Francisco
    Mica's AI agents replace the data ops teams fixing bad data. When bad or missing data breaks the pipeline, and orchestration, retries, and monitoring fail, painful manual review work kicks in, pulling humans in to investigate and patch data issues across systems. Mica does what those humans do: gathering the right information from internal docs and external systems, reasoning across context, and resolving errors autonomously to get the pipeline moving again. The result: dramatically reduce time, cost, and operational drag as your data pipelines scale without scaling ops headcount. Mica turns judgment-heavy data fixes from a manual bottleneck into an automated background process with full auditability.
    data-engineering
    data-science
    enterprise-software
  • kater.ai
    kater.ai
    Y Combinator LogoW2024
    Active • 3 employees • San Francisco
    1. You explain your problem. 2. Kater identifies the most important data questions to ask. 3. Kater writes the code. 4. You get insights in seconds rather than weeks. Kater.ai flips the script on enterprise analytics by making every user an expert analyst. It uses a continuous classification engine to turn a single business question into a contextualized package of questions that is specific to your needs. Kater puts the power of data into the hands of business experts while ensuring they use trusted data that is specific to their persona. No more waiting for data analysts. No more wasted time on analysis misfires and rework. Yvonne was a data engineer and analyst who built the entire data stack at CREXi. Robin led engineering in Microsoft. Data is the new oil. Companies are data-rich, insight-poor. We're helping companies become insight-rich. This is the future of data.
    data-engineering
    analytics
    artificial-intelligence
  • Reducto
    Reducto
    Y Combinator LogoW2024
    Active • 30 employees • San Francisco
    Reducto is the agentic document platform for leading AI teams building systems to automate document workflows at enterprise scale. Our platform provides a comprehensive toolkit for working with documents the way a human would, combining custom in-house and leading frontier models to power efficient and accurate document workflows. Reducto is trusted by leading AI teams at companies like Harvey, Scale AI, and Vanta. We are built for enterprise workloads with flexible deployment options from the cloud to fully air-gapped environments, SOC II and HIPAA compliance, and zero data retention. Learn more: https://reducto.ai/ Find our YC deal: https://reducto.ai/yc
    documents
    data-engineering
    search
    enterprise-software
    ai
  • DataShare
    DataShare
    Y Combinator LogoS2023
    Active • 1 employees • Austin, TX, USA
    DataShare is a data-as-a-service platform that lets you embed charts, dashboards and exports directly into your product. For example, if you run an accounting startup, DataShare would enable you to embed a full profit and loss dashboard, with downloadable statements. DataShare is backed by an enterprise-grade data warehouse, and can be implemented in fewer than 20 lines of code.
    data-engineering
    databases
    analytics
  • Cedalio
    Cedalio
    Y Combinator LogoS2023
    Active • 6 employees • San Francisco
    Cedalio is the AI agent platform for the Office of the CFO. We automate procurement, AP, and the data flows behind them — from invoice intake to 3-way match to payment — so finance teams stop doing the work and start governing it.
    data-engineering
    ai
    finance
    fintech
  • Artie
    Artie
    Y Combinator LogoS2023
    Active • 16 employees • San Francisco
    Artie is software that streams data from databases to data warehouses in real-time. Today, most companies run their ETL process every few hours or overnight, so their data warehouse is always out of date; with Artie, the warehouse always has live production data.
    data-engineering
    developer-tools
    enterprise-software
    saas
  • Ohm
    Ohm
    Y Combinator LogoW2023
    Active • 6 employees • San Francisco
    Ohm’s AI platform helps Fortune 100 engineering teams accelerate the engineering and testing of complex hardware, including wearables, electric vehicles, and batteries. Launched in 2025, Ohm's platform is used by leading teams to compress development cycles, optimize test programs, and launch better products, faster.
    data-engineering
    ai
    artificial-intelligence
    hardware
  • TableFlow
    TableFlow
    Y Combinator LogoW2023
    Active • 2 employees • San Francisco
    TableFlow builds AI teammates for data tasks, helping operations and data teams automate the messy, manual tasks buried in PDFs, spreadsheets, images, and emails.
    ai
    automation
    documents
    data-engineering
    saas
  • Honeydew
    Honeydew
    Y Combinator LogoW2023
    Active • 6 employees • Tel Aviv-Yafo, Israel
    The way people use data is constantly changing. Data teams must support every new context without breaking the shared truth. Honeydew’s semantic layer does it automatically. We validate each change and update every data flow. Using Honeydew, data teams can support 10x more data users - without more engineers or compromising integrity.
    saas
    data-engineering
    analytics
    b2b
  • Sunpia
    Sunpia
    Y Combinator LogoS2022
    Active • 3 employees • San Francisco
    Sunpia lets developers easily experience the cost and speed benefits of serverless infrastructure, without having to rewrite their code. Developers annotate their code and Sunpia automatically designs a microservice version of it they can deploy on their own cloud.
    data-engineering
    developer-tools
    kubernetes
  • Midplane
    Midplane
    Y Combinator LogoS2022
    Active • 2 employees • Berlin, Germany
    Companies are connecting AI coding agents like Cursor and Claude directly to their databases — where one wrong command can delete data or leak customer records. Midplane sits in between and controls what each agent is allowed to do: it blocks dangerous queries, hides sensitive data, and keeps a record of everything the agent touches. It’s open source, and runs either as a hosted service or on your own servers.
    b2b
    productivity
    generative-ai
    data-engineering
    ai
  • MovingLake
    MovingLake
    Y Combinator LogoS2022
    Active • 3 employees • Mexico City, CDMX, Mexico
    MovingLake is Fivetran for event-driven architectures. Companies such as Casai use our product to obtain orders and price changes in real time.
    analytics
    api
    data-engineering
    b2b
    saas
  • Lamin
    Lamin
    Y Combinator LogoS2022
    Active • 10 employees
    Improve your workflows with open-source context and data access. Query, trace, and govern with a lineage-native, format-agnostic lakehouse for agents and teams. With support for biological formats and registries — by the creators of Scanpy
    biotech
    data-engineering
    machine-learning
    developer-tools
    open-source
  • Sanifu
    Sanifu
    Y Combinator LogoS2022
    Active • 5 employees • Nairobi, Kenya
    Sanifu automates the repetitive work spreadsheets are used for. Teams across finance, marketing, operations, and revenue simply describe their repetitive work like cleaning and validating data, matching records, preparing reports and moving data from emails, PDFs and Excel exports into dashboards and other systems. Sanifu captures the steps, rules and outputs so the work is repeated reliably every time Some of these work never needed spreadsheets to begin with; the likes of data cleanup, bank reconciliations, report consolidation, payment allocations, channel attributions. Spreadsheets have become the standard workbench for handling these tasks and in the process, teams lose valuable time doing manual data entry, filling templates, formatting data, applying formulas, deleting rows or columns. Sanifu is replacing the spreadsheet while maintaining the logic and processes behind the work, getting you from inputs to your desired outputs in seconds. This significantly reduces time spent on copy-paste work, spreadsheet cleanup, and manual entries and checking, while improving accuracy, consistency, and operational visibility. Sanifu helps teams scale critical back-office processes without relying on brittle spreadsheets, scripts, or general AI tools that reinvent the logic every run. Sanifu is a full operational platform with 5k+ app integrations, 1k+ native functions, forms, scheduler, data stores, dashboards, approvals, reviews, AI-powered entity matching among many other features that make it an end-to-end data app management platform
    b2b
    saas
    data-engineering
    analytics
    ai
  • LanceDB
    LanceDB
    Y Combinator LogoW2022
    Active • 35 employees • San Francisco
    LanceDB is a new open-source vector database that can support low-latency billion-scale vector search on a single node. Built around a new columnar data format, LanceDB makes it incredibly easy to build applications for generative AI, recsys, search engines, content moderation, and more.
    aiops
    data-engineering
    machine-learning
    open-source
  • Elementary
    Elementary
    Y Combinator LogoW2022
    Active • 12 employees • Tel Aviv-Yafo, Israel
    Elementary enables data teams to detect problems in their data before their users do. An open-source solution that any data engineer can deploy in minutes without sharing sensitive data.
    data-engineering
    analytics
    developer-tools
    open-source
  • Dynamo AI
    Dynamo AI
    Y Combinator LogoW2022
    Active • 40 employees • San Francisco
    End-to-end privacy, security, and compliance solutions to prepare your organization for emerging AI regulations.
    privacy
    data-engineering
    machine-learning
  • Sieve
    Sieve
    Y Combinator LogoW2022
    Active • 34 employees • San Francisco
    Sieve builds the data and environments frontier AI labs use to train the next generation of multimodal systems. AI is moving beyond chatbots into video, audio, images, software, robotics, and interactive worlds. The next generation of models will need to understand how the world looks, sounds, moves, responds, and changes over time. Progress is bottlenecked by one thing: high-quality data. Sieve brings together exabyte-scale infrastructure, novel multimodal understanding techniques, large-scale sourcing, and deep research partnerships to create datasets and environments with unmatched precision, quality, and speed. This has earned the trust of frontier AI labs, Fortune 100 companies, and fast-growing AI startups working on generative media, robotics, computer use, world models, and agentic systems.
    video
    developer-tools
    data-engineering
    data-labeling
    ai
  • Versable
    Versable
    Y Combinator LogoW2022
    Active • 3 employees • San Francisco
    Auto parts retailers get product data from hundreds of manufacturers that is inaccurate and inconsistent, often with big gaps in key values. They currently have a team of "catalog managers" who are required to process and enhance this data line by line, resulting in a week to months long lag between receiving product data and actually being able to start generating revenue from those products. Versable leverages AI to scan the web for tens of millions of auto parts listings, and uses a fine-tuned LLM with RAG to instantly process, enhance, and transform data. With just a part number, Versable is able to generate market-ready titles, product descriptions, and specs, in any format that's needed.
    automotive
    manufacturing
    data-engineering
    ai
  • Trackingplan
    Trackingplan
    Y Combinator LogoW2022
    Active • 17 employees • Barcelona, Spain
    Trackingplan automatically discovers and monitors all the information your applications and websites are collecting, ensuring that you can trust your BI, analytics, marketing, and sales tools. You can think of us as Segment Protocols but totally transparent, where developers can keep using Google Analytics, Amplitude, Hubspot, Intercom, Braze, etc. as they are used to. Installed in minutes in using your Tag Manager or adding just one line of code to your web or apps, we model all the data being sent to third parties. Since Trackingplan understands what each piece of data means, it identifies patterns, detects anomalies, and automatically connects the dots to create value from data that was hidden in plain sight: - An always up-to-date single source of truth and data governance tool. To discover, understand and document your data and improve communication across teams. - Automated notifications when something breaks or changes. To make sure that integrations are always well implemented: Schema errors, traffic anomalies, rogue events... - Easy to understand, customizable, cross-service alerts. To detect trends, insights, and problems without using complex, engineer-oriented solutions.
    data-engineering
    analytics
    saas
  • Pipekit
    Pipekit
    Y Combinator LogoS2021
    Active • 9 employees
    Our app manages Argo Workflows for data teams, enabling complex data & CI pipelines in half the time while saving companies hundreds of thousands of dollars annually. We maintain Argo Workflows, an open-source pipeline framework for Kubernetes that’s used in production by Bloomberg, Intuit, Adobe, New Relic, NVIDIA, and many other open-source early adopters.
    open-source
    developer-tools
    data-engineering
    devops
  • Evidence
    Evidence
    Y Combinator LogoS2021
    Active • 6 employees • Toronto, ON, Canada
    Evidence is an open source, code-based alternative to drag-and-drop BI tools. Build polished data products with just SQL and markdown.
    b2b
    data-engineering
    developer-tools
    analytics
    data-visualization
  • Whaly
    Whaly
    Y Combinator LogoS2021
    Active • 3 employees • Paris, France
    Whaly helps data teams save time on maintenance and analysis building while making business users more autonomous on the analysis they want to improve their decision making. We do this by providing a self service data platform where both data and business teams can work together. We understood that most data teams were ending up being a bottleneck for the rest of the company and needed to give more autonomy to business teams to back their decisions with data. Emilien, Florian and Pierre were the minds behind the Data advertising platforms of the major media and e-commerce companies in France in their earlier position as Product Manager and head of Customer Success, giving them an edge on how to execute successfully a data project.
    data-engineering
Loading more companies...